Papers with discriminative formulation
Retrieving Multimodal Prompts for Generative Visual Question Answering (2023.findings-acl)
Copied to clipboard
| Challenge: | Visual question answering (VQA) is a multimodal machine learning problem that challenges a model to answer a question posed about an image. |
| Approach: | They propose a generative model enhanced by multimodal prompt retrieval that integrates retrieved prompts and multimodal features to generate answers in free text. |
| Outcome: | The proposed model outperforms its non-retrieval counterpart by 30% on medical VQA tasks. |
Radical Allomorphy: Phonological Surface Forms without Phonology (2025.findings-emnlp)
Copied to clipboard
| Challenge: | Recent work typically frames morphophonology as generating surface forms from abstract underlying representations (URs) this theory-laden assumption is expensive to annotate, especially in low-resource settings. |
| Approach: | a new approach frames morphophonology as generating surface forms from abstract underlying representations by applying phonological rules or constraints. |
| Outcome: | The proposed model removes the need to posit or label URs and lets the model exploit the surface evidence directly. |